Enhanced statistical rankings via targeted data collection

نویسندگان

  • Braxton Osting
  • Christoph Brune
  • Stanley Osher
چکیده

Given a graph where vertices represent alternatives and pairwise comparison data, yij , is given on the edges, the statistical ranking problem is to find a potential function, defined on the vertices, such that the gradient of the potential function agrees with pairwise comparisons. We study the dependence of the statistical ranking problem on the available pairwise data, i.e., pairs (i, j) for which the pairwise comparison data yij is known, and propose a framework to identify data which, when augmented with the current dataset, maximally increases the Fisher information of the ranking. Under certain assumptions, the data collection problem decouples, reducing to a problem of finding an edge set on the graph (with a fixed number of edges) such that the second eigenvalue of the graph Laplacian is maximal. This reduction of the data collection problem to a spectral graph-theoretic question is one of the primary contributions of this work. As an application, we study the Yahoo! Movie user rating dataset and demonstrate that the addition of a small number of well-chosen pairwise comparisons can significantly increase the Fisher informativeness of the ranking.

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Optimal Data Collection for Improved Rankings Expose Well-Connected Graphs

Given a graph where vertices represent alternatives and arcs represent pairwise comparison data, the statistical ranking problem is to find a potential function, defined on the vertices, such that the gradient of the potential function agrees with the pairwise comparisons. Our goal in this paper is to develop a method for collecting data for which the least squares estimator for the ranking pro...

متن کامل

A Bi-Functional Targeted P28-NRC Chimeric Protein with Enhanced Cytotoxic Effects on Breast Cancer Cell Lines

One of the emerging therapeutic strategies for targeted therapy of cancer is the use of chimeric proteins. The p28 peptide has the ability of selective entrance and activating apoptosis in breast cancer cells. The NRC antimicrobial peptide showed cytotoxic activity on various breast cancer including drug-resistant variants but also normal cell lines. Here we designed a chimeric protein consiste...

متن کامل

A Bi-Functional Targeted P28-NRC Chimeric Protein with Enhanced Cytotoxic Effects on Breast Cancer Cell Lines

One of the emerging therapeutic strategies for targeted therapy of cancer is the use of chimeric proteins. The p28 peptide has the ability of selective entrance and activating apoptosis in breast cancer cells. The NRC antimicrobial peptide showed cytotoxic activity on various breast cancer including drug-resistant variants but also normal cell lines. Here we designed a chimeric protein consiste...

متن کامل

Optimal data collection for informative rankings expose well-connected graphs

Given a graph where vertices represent alternatives and arcs represent pairwise comparison data, the statistical ranking problem is to find a potential function, defined on the vertices, such that the gradient of the potential function agrees with the pairwise comparisons. Our goal in this paper is to develop a method for collecting data for which the least squares estimator for the ranking pro...

متن کامل

What is the effect of country-specific characteristics on the research performance of scientific institutions? Using multi-level statistical models to rank and map universities and research-focused institutions worldwide

Bornmann, Stefaner, de Moya Anegón, and Mutz (in press) have introduced a web application (www.excellencemapping.net) which is linked to both academic ranking lists published hitherto (e.g. the Academic Ranking of World Universities) as well as spatial visualization approaches. The web application visualizes institutional performance within specific subject areas as ranking lists and on custom ...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 2013